Papers by Zain Muhammad Mujahid
Qorǵau: Evaluating Safety in Kazakh-Russian Bilingual Contexts (2025.findings-acl)
Copied to clipboard
Maiya Goloburda, Nurkhan Laiyk, Diana Turmakhan, Yuxia Wang, Mukhammed Togmanov, Jonibek Mansurov, Askhat Sametov, Nurdaulet Mukhituly, Minghan Wang, Daniil Orel, Zain Muhammad Mujahid, Fajri Koto, Timothy Baldwin, Preslav Nakov
| Challenge: | Large language models (LLMs) have the potential to generate harmful content, posing risks to users. |
| Approach: | They propose a dataset specifically designed for safety evaluation in Kazakh and Russian . they use a bilingual context in Kazakhstan where both Kazakh (a low-resource language) and Russian (a high-resourced language) |
| Outcome: | The proposed dataset is designed for safety evaluation in Kazakh and Russian . it shows that both multilingual and language-specific LLMs perform better than others . |
Unstructured Evidence Attribution for Long Context Query Focused Summarization (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing systems struggle to copy and properly cite unstructured evidence, which also tends to be “lost-in-the-middle”. |
| Approach: | They propose to extract unstructured evidence spans to improve the trustworthiness of large language models by citing unstructure . they propose to use this dataset as a training supervision for unstructure-based evidence summarization. |
| Outcome: | The proposed pipeline generates more relevant and factually consistent evidence than baselines with no fine-tuning and fixed granularity evidence. |
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis (2026.findings-acl)
Copied to clipboard
Muhammad Arslan Manzoor, Dilshod Azizov, Daniil Orel, Umer Siddique, Zain Muhammad Mujahid, Yufang Hou, Preslav Nakov
| Challenge: | a large-scale label set for media outlets from Media Bias/Fact Check (MBFC) is lacking in the field. |
| Approach: | They propose to use a large-scale label set to analyze outlets' representations . they also propose to evaluate embedding views and fusion strategies . |
| Outcome: | The proposed method achieves state-of-the-art results on ACL-2020 and establishes strong benchmarks on MBFC-2025. |
Stress Testing Factual Consistency Metrics for Long-Document Summarization (2026.acl-long)
Copied to clipboard
| Challenge: | Existing short-form summarization metrics struggle with input length limitations and long-range dependencies. |
| Approach: | They propose to evaluate the reliability of six widely used reference-free factuality metrics in the long-document setting by applying seven factually-preserving perturbations to summaries. |
| Outcome: | The proposed short-form summarization metrics struggle with long-range dependencies and input length limitations. |
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts (2025.findings-acl)
Copied to clipboard
| Challenge: | Important efforts to characterize news media outlets in terms of their political bias and factuality are labor-intensive and prone to human biases. |
| Approach: | They propose a method that emulates criteria used by professional fact-checkers to assess the factuality and political bias of an entire outlet. |
| Outcome: | The proposed method improves on baselines and with multiple LLMs. |